Papers with cultural grounding

3 papers
FFE-Hallu: Hallucinations in Fixed Figurative Expressions: A Benchmark of Idioms and Proverbs in the Persian Language (2026.eacl-long)

Copied to clipboard

Challenge: Figurative language, especially fixed figurative expressions, poses unique challenges for large language models . Unlike literal phrases, FFEs are culturally grounded and often non-compositional, making them vulnerable to figurativ hallucination .
Approach: They propose a benchmark to evaluate LLMs' ability to generate, detect, and translate fixed figurative expressions in Persian.
Outcome: The proposed benchmarks show that LLMs still struggle with figurative language expressions . the benchmarks are based on 600 carefully curated examples spanning three tasks .
Pearl: A Multimodal Culturally-Aware Arabic Instruction Dataset (2025.findings-emnlp)

Copied to clipboard

Challenge: Mainstream large vision-language models (LVLMs) inherently encode cultural biases, highlighting the need for diverse multimodal datasets.
Approach: They propose to construct a large-scale Arabic multimodal dataset and benchmark explicitly designed for cultural understanding.
Outcome: The proposed dataset covers ten culturally significant domains covering all Arab countries and includes two evaluation benchmarks (PEARL and PEARL-LITE) and a specialized subset (PearL-X).
When Cultures Meet: Multicultural Text-to-Image Generation (2026.findings-acl)

Copied to clipboard

Challenge: a new task to evaluate text-to-image generation models for multicultural scenes is unexplored.
Approach: They propose a benchmark task to evaluate text-to-image models in multicultural settings . they use a dataset of 9,000 images spanning five countries, three age groups, two genders, 25 historical landmarks, and five languages to analyze behavior .
Outcome: The proposed benchmark analyzes the behavior of state-of-the-art models across multiple dimensions including alignment, image quality, aesthetics, knowledge, and fairness.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations